Papers by Shei Pern Chua

1 papers
Between a Rock and a Hard Place: The Tension Between Ethical Reasoning and Safety Alignment in LLMs (2026.acl-long)

Copied to clipboard

Challenge: Large Language Models (LLMs) safety alignment predominantly operates on a binary assumption that requests are either safe or unsafe.
Approach: They propose a methodology that embeds harmful requests within ethical framings to exploit this vulnerability.
Outcome: The proposed framework achieves high success rates by exploiting model's own ethical reasoning to frame harmful actions as morally necessary compromises.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations